NSF PAR Search | NSF Public Access Repository

Note: When clicking on a Digital Object Identifier (DOI) number, you will be taken to an external site maintained by the publisher. Some full text articles may not yet be available without a charge during the embargo (administrative interval).
What is a DOI Number?

Some links on this page may take you to non-federal websites. Their policies may differ from this site.

Using Developer Discussions to Guide Fixing Bugs in Software

Panthaplackel, Sheena; Gligoric, Milos; Li, Junyi Jessy; Mooney, Raymond (December 2022, Findings of the Association for Computational Linguistics: EMNLP 2022)

Automatically fixing software bugs is a challenging task. While recent work showed that natural language context is useful in guiding bug-fixing models, the approach required prompting developers to provide this context, which was simulated through commit messages written after the bug-fixing code changes were made. We instead propose using bug report discussions, which are available before the task is performed and are also naturally occurring, avoiding the need for any additional information from developers. For this, we augment standard bug-fixing datasets with bug report discussions. Using these newly compiled datasets, we demonstrate that various forms of natural language context derived from such discussions can aid bug-fixing, even leading to improved performance over using commit messages corresponding to the oracle bug-fixing commits.
more » « less
Full Text Available
CoditT5: Pretraining for Source Code and Natural Language Editing

https://doi.org/10.1145/3551349.3556955

Zhang, Jiyang; Panthaplackel, Sheena; Nie, Pengyu; Li, Junyi Jessy; Gligoric, Milos (October 2022, CoditT5: Pretraining for Source Code and Natural Language Editing)

Full Text Available
Learning to Describe Solutions for Bug Reports Based on Developer Discussions

https://doi.org/10.18653/v1/2022.findings-acl.231

Panthaplackel, Sheena; Li, Junyi Jessy; Gligoric, Milos; Mooney, Ray (January 2022, Findings of the Association for Computational Linguistics: ACL 2022)

When a software bug is reported, developers engage in a discussion to collaboratively resolve it. While the solution is likely formulated within the discussion, it is often buried in a large amount of text, making it difficult to comprehend and delaying its implementation. To expedite bug resolution, we propose generating a concise natural language description of the solution by synthesizing relevant content within the discussion, which encompasses both natural language and source code. We build a corpus for this task using a novel technique for obtaining noisy supervision from repository changes linked to bug reports, with which we establish benchmarks. We also design two systems for generating a description during an ongoing discussion by classifying when sufficient context for performing the task emerges in real-time. With automated and human evaluation, we find this task to form an ideal testbed for complex reasoning in long, bimodal dialogue context.
more » « less
Full Text Available
Associating Natural Language Comment and Source Code Entities

https://doi.org/10.1609/aaai.v34i05.6382

Panthaplackel, Sheena; Gligoric, Milos; Mooney, Raymond J.; Li, Junyi Jessy (April 2020, Proceedings of the AAAI Conference on Artificial Intelligence)

Comments are an integral part of software development; they are natural language descriptions associated with source code elements. Understanding explicit associations can be useful in improving code comprehensibility and maintaining the consistency between code and comments. As an initial step towards this larger goal, we address the task of associating entities in Javadoc comments with elements in Java source code. We propose an approach for automatically extracting supervised data using revision histories of open source projects and present a manually annotated evaluation dataset for this task. We develop a binary classifier and a sequence labeling model by crafting a rich feature set which encompasses various aspects of code, comments, and the relationships between them. Experiments show that our systems outperform several baselines learning from the proposed supervision.
more » « less
Full Text Available
Learning to Update Natural Language Comments Based on Code Changes

https://doi.org/10.18653/v1/2020.acl-main.168

Panthaplackel, Sheena; Nie, Pengyu; Gligoric, Milos; Li, Junyi Jessy; Mooney, Raymond (January 2020, Association for Computational Linguistics)
null (Ed.)
Full Text Available

Search for: All records